Papers with robust detectors

2 papers
On the Zero-Shot Generalization of Machine-Generated Text Detectors (2023.findings-emnlp)

Copied to clipboard

Challenge: rampant proliferation of large language models generates text indistinguishable from human-written language.
Approach: They train neural detectors on outputs of a new generator and test their performance on held-out generators.
Outcome: The proposed detectors can be built on training data from medium-sized models.
RAFT: Realistic Attacks to Fool Text Detectors (2024.emnlp-main)

Copied to clipboard

Challenge: Large language models (LLMs) have exhibited remarkable fluency across tasks, but their unethical applications are unclear.
Approach: They propose a grammar error-free black-box attack that exploits LLM embeddings at the word-level while preserving original text quality.
Outcome: The proposed attack compromises all detectors across domains and is transferable across source models.

What is GenGO?

GenGO is an NLP powered publication search system. It currenctly indexes 30k+ papers from ACL Anthology, and implements multi-aspect summarization, semantic search, and more!

Information

About
Limitations